The Register -
15 Jul 2024 15:01
They're gonna need a heck of a lot of memory bandwidth - not to mention capacity - to do it Analysis Within a few years, AMD expects to have notebook chips capable of running 30 billion parameter large language models locally at a speedy 100 tokens per second....
Share this Article